UniChem: a unified chemical structure cross-referencing and identifier tracking system

نویسندگان

  • Jon Chambers
  • Mark Davies
  • Anna Gaulton
  • Anne Hersey
  • Sameer Velankar
  • Robert Petryszak
  • Janna Hastings
  • Louisa J. Bellis
  • Shaun McGlinchey
  • John P. Overington
چکیده

UniChem is a freely available compound identifier mapping service on the internet, designed to optimize the efficiency with which structure-based hyperlinks may be built and maintained between chemistry-based resources. In the past, the creation and maintenance of such links at EMBL-EBI, where several chemistry-based resources exist, has required independent efforts by each of the separate teams. These efforts were complicated by the different data models, release schedules, and differing business rules for compound normalization and identifier nomenclature that exist across the organization. UniChem, a large-scale, non-redundant database of Standard InChIs with pointers between these structures and chemical identifiers from all the separate chemistry resources, was developed as a means of efficiently sharing the maintenance overhead of creating these links. Thus, for each source represented in UniChem, all links to and from all other sources are automatically calculated and immediately available for all to use. Updated mappings are immediately available upon loading of new data releases from the sources. Web services in UniChem provide users with a single simple automatable mechanism for maintaining all links from their resource to all other sources represented in UniChem. In addition, functionality to track changes in identifier usage allows users to monitor which identifiers are current, and which are obsolete. Lastly, UniChem has been deliberately designed to allow additional resources to be included with minimal effort. Indeed, the recent inclusion of data sources external to EMBL-EBI has provided a simple means of providing users with an even wider selection of resources with which to link to, all at no extra cost, while at the same time providing a simple mechanism for external resources to link to all EMBL-EBI chemistry resources.

برای دانلود رایگان متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

UniChem: Quick tour

UniChem is a simple, large-scale, non-redundant database of pointers between chemical structure identifiers. Its purpose is to optimise the efficiency with which structure-based hyperlinks are built and maintained between chemistry databases. This is particularly suitable for creating links 'on the fly' by the use of REST web services. Primarily, this service has been designed to maintain cross...

متن کامل

Iterative learning identification and control for dynamic systems described by NARMAX model

A new iterative learning controller is proposed for a general unknown discrete time-varying nonlinear non-affine system represented by NARMAX (Nonlinear Autoregressive Moving Average with eXogenous inputs) model. The proposed controller is composed of an iterative learning neural identifier and an iterative learning controller. Iterative learning control and iterative learning identification ar...

متن کامل

UniChem: extension of InChI-based compound mapping to salt, connectivity and stereochemistry layers

UniChem is a low-maintenance, fast and freely available compound identifier mapping service, recently made available on the Internet. Until now, the criterion of molecular equivalence within UniChem has been on the basis of complete identity between Standard InChIs. However, a limitation of this approach is that stereoisomers, isotopes and salts of otherwise identical molecules are not consider...

متن کامل

طرح ادغام سرشاخه خوشه طب سنتی ایران در ساختار ابَراصطلاحنامه « نظام زبان واحد پزشکی (UMLS)»

Background & Aim: Unified Medical Language System (UMLS) is an extensive ontology of biomedical knowledge developed and maintained by U.S. National Library of Medicine (NLM). Traditional Iranian Medicine (TIM) does not have any position in the structure of metathesaurus of UMLS. The main aim of this study was designing a scheme of TIM cluster's crotch mapping in the structure of metathesaurus o...

متن کامل

Unique identifiers for small molecules enable rigorous labeling of their atoms

Rigorous characterization of small organic molecules in terms of their structural and biological properties is vital to biomedical research. The three-dimensional structure of a molecule, its 'photo ID', is inefficient for searching and matching tasks. Instead, identifiers play a key role in accessing compound data. Unique and reproducible molecule and atom identifiers are required to ensure th...

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

عنوان ژورنال:

دوره 5  شماره 

صفحات  -

تاریخ انتشار 2013